PROJET - Spatial Audio Separation Using Projections
Identifieur interne : 000169 ( Main/Exploration ); précédent : 000168; suivant : 000170PROJET - Spatial Audio Separation Using Projections
Auteurs : Derry Fitzgerald [Irlande (pays)] ; Antoine Liutkus [France] ; Roland Badeau [France]Source :
English descriptors
Abstract
We propose a projection-based method for the unmixing of multi-channel audio signals into their different constituent spatial objects. Here, spatial objects are modelled using a unified framework which handles both point sources and diffuse sources. We then propose a novel methodology to estimate and take advantage of the spatial dependencies of an object. Where previous research has processed the original multichannel mixtures directly and has been principally focused on the use of inter-channel covariance structures, here we instead process projections of the multichannel signal on many different spatial directions. These linear combinations consist of observations where some spatial objects are cancelled or enhanced. We then propose an algorithm which takes these projections as the observations, discarding dependencies between them. Since each one contains global information regarding all channels of the original multichannel mixture, this provides an effective means of learning the parameters of the original audio, while avoiding the need for joint-processing of all the channels. We further show how to recover the separated spatial objects and demonstrate the use of the technique on stereophonic music signals.
Url:
Affiliations:
Links toward previous steps (curation, corpus...)
- to stream Hal, to step Corpus: 003A58
- to stream Hal, to step Curation: 003A58
- to stream Hal, to step Checkpoint: 000120
- to stream Main, to step Merge: 000162
- to stream Main, to step Curation: 000169
Le document en format XML
<record><TEI><teiHeader><fileDesc><titleStmt><title xml:lang="en">PROJET - Spatial Audio Separation Using Projections</title>
<author><name sortKey="Fitzgerald, Derry" sort="Fitzgerald, Derry" uniqKey="Fitzgerald D" first="Derry" last="Fitzgerald">Derry Fitzgerald</name>
<affiliation wicri:level="1"><hal:affiliation type="laboratory" xml:id="struct-246988" status="VALID"><orgName>NIMBUS Centre [Cork]</orgName>
<orgName type="acronym">NIMBUS</orgName>
<desc><address><addrLine>Cork Institute of Technology Campus, Bishopstown, Cork</addrLine>
<country key="IE"></country>
</address>
<ref type="url">http://nimbus.cit.ie</ref>
</desc>
<listRelation><relation active="#struct-313560" type="direct"></relation>
</listRelation>
<tutelles><tutelle active="#struct-313560" type="direct"><org type="institution" xml:id="struct-313560" status="INCOMING"><orgName>Cork Institute of Technology</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>Irlande (pays)</country>
</affiliation>
</author>
<author><name sortKey="Liutkus, Antoine" sort="Liutkus, Antoine" uniqKey="Liutkus A" first="Antoine" last="Liutkus">Antoine Liutkus</name>
<affiliation wicri:level="1"><hal:affiliation type="researchteam" xml:id="struct-420403" status="VALID"><idno type="RNSR">201421147E</idno>
<orgName>Speech Modeling for Facilitating Oral-Based Communication</orgName>
<orgName type="acronym">MULTISPEECH</orgName>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/equipes/multispeech</ref>
</desc>
<listRelation><relation active="#struct-129671" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-423086" type="direct"></relation>
<relation active="#struct-206040" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
<tutelles><tutelle active="#struct-129671" type="direct"><org type="laboratory" xml:id="struct-129671" status="VALID"><idno type="RNSR">198618246Y</idno>
<orgName>INRIA Nancy - Grand Est</orgName>
<desc><address><addrLine>615 rue du Jardin Botanique 54600 Villers-lès-Nancy</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/nancy</ref>
</desc>
<listRelation><relation active="#struct-300009" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-300009" type="indirect"><org type="institution" xml:id="struct-300009" status="VALID"><orgName>Institut National de Recherche en Informatique et en Automatique</orgName>
<orgName type="acronym">Inria</orgName>
<desc><address><addrLine>Domaine de VoluceauRocquencourt - BP 10578153 Le Chesnay Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/en/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-423086" type="direct"><org type="department" xml:id="struct-423086" status="VALID"><orgName>Department of Natural Language Processing & Knowledge Discovery</orgName>
<orgName type="acronym">LORIA - NLPKD</orgName>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr/la-recherche-en/departements/Knowledge-and-Language-Management</ref>
</desc>
<listRelation><relation active="#struct-206040" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-206040" type="indirect"><org type="laboratory" xml:id="struct-206040" status="VALID"><idno type="IdRef">067077927</idno>
<idno type="RNSR">198912571S</idno>
<idno type="IdUnivLorraine">[UL]RSI--</idno>
<orgName>Laboratoire Lorrain de Recherche en Informatique et ses Applications</orgName>
<orgName type="acronym">LORIA</orgName>
<date type="start">2012-01-01</date>
<desc><address><addrLine>Campus Scientifique BP 239 54506 Vandoeuvre-lès-Nancy Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr</ref>
</desc>
<listRelation><relation active="#struct-300009" type="direct"></relation>
<relation active="#struct-413289" type="direct"></relation>
<relation name="UMR7503" active="#struct-441569" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-413289" type="indirect"><org type="institution" xml:id="struct-413289" status="VALID"><idno type="IdRef">157040569</idno>
<idno type="IdUnivLorraine">[UL]100--</idno>
<orgName>Université de Lorraine</orgName>
<orgName type="acronym">UL</orgName>
<date type="start">2012-01-01</date>
<desc><address><addrLine>34 cours Léopold - CS 25233 - 54052 Nancy cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.univ-lorraine.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR7503" active="#struct-441569" type="indirect"><org type="institution" xml:id="struct-441569" status="VALID"><idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
<placeName><settlement type="city">Nancy</settlement>
<settlement type="city">Metz</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Université de Lorraine</orgName>
</affiliation>
</author>
<author><name sortKey="Badeau, Roland" sort="Badeau, Roland" uniqKey="Badeau R" first="Roland" last="Badeau">Roland Badeau</name>
<affiliation wicri:level="1"><hal:affiliation type="laboratory" xml:id="struct-162010" status="VALID"><orgName>Laboratoire Traitement et Communication de l'Information</orgName>
<orgName type="acronym">LTCI</orgName>
<desc><address><addrLine>46 rue Barrault F-75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.ltci.telecom-paristech.fr/</ref>
</desc>
<listRelation><relation active="#struct-300362" type="direct"></relation>
<relation active="#struct-302102" type="direct"></relation>
<relation name="UMR5141" active="#struct-441569" type="direct"></relation>
</listRelation>
<tutelles><tutelle active="#struct-300362" type="direct"><org type="institution" xml:id="struct-300362" status="VALID"><orgName>Télécom ParisTech</orgName>
<desc><address><addrLine>46 rue Barrault 75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.telecom-paristech.fr</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-302102" type="direct"><org type="institution" xml:id="struct-302102" status="VALID"><orgName>Institut Mines-Télécom</orgName>
<desc><address><addrLine>46 rue Barrault -75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.mines-telecom.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR5141" active="#struct-441569" type="direct"><org type="institution" xml:id="struct-441569" status="VALID"><idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
</titleStmt>
<publicationStmt><idno type="wicri:source">HAL</idno>
<idno type="RBID">Hal:hal-01248014</idno>
<idno type="halId">hal-01248014</idno>
<idno type="halUri">https://hal.archives-ouvertes.fr/hal-01248014</idno>
<idno type="url">https://hal.archives-ouvertes.fr/hal-01248014</idno>
<date when="2016">2016</date>
<idno type="wicri:Area/Hal/Corpus">003A58</idno>
<idno type="wicri:Area/Hal/Curation">003A58</idno>
<idno type="wicri:Area/Hal/Checkpoint">000120</idno>
<idno type="wicri:explorRef" wicri:stream="Hal" wicri:step="Checkpoint">000120</idno>
<idno type="wicri:Area/Main/Merge">000162</idno>
<idno type="wicri:Area/Main/Curation">000169</idno>
<idno type="wicri:Area/Main/Exploration">000169</idno>
</publicationStmt>
<sourceDesc><biblStruct><analytic><title xml:lang="en">PROJET - Spatial Audio Separation Using Projections</title>
<author><name sortKey="Fitzgerald, Derry" sort="Fitzgerald, Derry" uniqKey="Fitzgerald D" first="Derry" last="Fitzgerald">Derry Fitzgerald</name>
<affiliation wicri:level="1"><hal:affiliation type="laboratory" xml:id="struct-246988" status="VALID"><orgName>NIMBUS Centre [Cork]</orgName>
<orgName type="acronym">NIMBUS</orgName>
<desc><address><addrLine>Cork Institute of Technology Campus, Bishopstown, Cork</addrLine>
<country key="IE"></country>
</address>
<ref type="url">http://nimbus.cit.ie</ref>
</desc>
<listRelation><relation active="#struct-313560" type="direct"></relation>
</listRelation>
<tutelles><tutelle active="#struct-313560" type="direct"><org type="institution" xml:id="struct-313560" status="INCOMING"><orgName>Cork Institute of Technology</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>Irlande (pays)</country>
</affiliation>
</author>
<author><name sortKey="Liutkus, Antoine" sort="Liutkus, Antoine" uniqKey="Liutkus A" first="Antoine" last="Liutkus">Antoine Liutkus</name>
<affiliation wicri:level="1"><hal:affiliation type="researchteam" xml:id="struct-420403" status="VALID"><idno type="RNSR">201421147E</idno>
<orgName>Speech Modeling for Facilitating Oral-Based Communication</orgName>
<orgName type="acronym">MULTISPEECH</orgName>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/equipes/multispeech</ref>
</desc>
<listRelation><relation active="#struct-129671" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-423086" type="direct"></relation>
<relation active="#struct-206040" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
<tutelles><tutelle active="#struct-129671" type="direct"><org type="laboratory" xml:id="struct-129671" status="VALID"><idno type="RNSR">198618246Y</idno>
<orgName>INRIA Nancy - Grand Est</orgName>
<desc><address><addrLine>615 rue du Jardin Botanique 54600 Villers-lès-Nancy</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/nancy</ref>
</desc>
<listRelation><relation active="#struct-300009" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-300009" type="indirect"><org type="institution" xml:id="struct-300009" status="VALID"><orgName>Institut National de Recherche en Informatique et en Automatique</orgName>
<orgName type="acronym">Inria</orgName>
<desc><address><addrLine>Domaine de VoluceauRocquencourt - BP 10578153 Le Chesnay Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/en/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-423086" type="direct"><org type="department" xml:id="struct-423086" status="VALID"><orgName>Department of Natural Language Processing & Knowledge Discovery</orgName>
<orgName type="acronym">LORIA - NLPKD</orgName>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr/la-recherche-en/departements/Knowledge-and-Language-Management</ref>
</desc>
<listRelation><relation active="#struct-206040" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-206040" type="indirect"><org type="laboratory" xml:id="struct-206040" status="VALID"><idno type="IdRef">067077927</idno>
<idno type="RNSR">198912571S</idno>
<idno type="IdUnivLorraine">[UL]RSI--</idno>
<orgName>Laboratoire Lorrain de Recherche en Informatique et ses Applications</orgName>
<orgName type="acronym">LORIA</orgName>
<date type="start">2012-01-01</date>
<desc><address><addrLine>Campus Scientifique BP 239 54506 Vandoeuvre-lès-Nancy Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr</ref>
</desc>
<listRelation><relation active="#struct-300009" type="direct"></relation>
<relation active="#struct-413289" type="direct"></relation>
<relation name="UMR7503" active="#struct-441569" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-413289" type="indirect"><org type="institution" xml:id="struct-413289" status="VALID"><idno type="IdRef">157040569</idno>
<idno type="IdUnivLorraine">[UL]100--</idno>
<orgName>Université de Lorraine</orgName>
<orgName type="acronym">UL</orgName>
<date type="start">2012-01-01</date>
<desc><address><addrLine>34 cours Léopold - CS 25233 - 54052 Nancy cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.univ-lorraine.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR7503" active="#struct-441569" type="indirect"><org type="institution" xml:id="struct-441569" status="VALID"><idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
<placeName><settlement type="city">Nancy</settlement>
<settlement type="city">Metz</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Université de Lorraine</orgName>
</affiliation>
</author>
<author><name sortKey="Badeau, Roland" sort="Badeau, Roland" uniqKey="Badeau R" first="Roland" last="Badeau">Roland Badeau</name>
<affiliation wicri:level="1"><hal:affiliation type="laboratory" xml:id="struct-162010" status="VALID"><orgName>Laboratoire Traitement et Communication de l'Information</orgName>
<orgName type="acronym">LTCI</orgName>
<desc><address><addrLine>46 rue Barrault F-75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.ltci.telecom-paristech.fr/</ref>
</desc>
<listRelation><relation active="#struct-300362" type="direct"></relation>
<relation active="#struct-302102" type="direct"></relation>
<relation name="UMR5141" active="#struct-441569" type="direct"></relation>
</listRelation>
<tutelles><tutelle active="#struct-300362" type="direct"><org type="institution" xml:id="struct-300362" status="VALID"><orgName>Télécom ParisTech</orgName>
<desc><address><addrLine>46 rue Barrault 75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.telecom-paristech.fr</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-302102" type="direct"><org type="institution" xml:id="struct-302102" status="VALID"><orgName>Institut Mines-Télécom</orgName>
<desc><address><addrLine>46 rue Barrault -75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.mines-telecom.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR5141" active="#struct-441569" type="direct"><org type="institution" xml:id="struct-441569" status="VALID"><idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
</analytic>
</biblStruct>
</sourceDesc>
</fileDesc>
<profileDesc><textClass><keywords scheme="mix" xml:lang="en"><term>Spatial Projection</term>
<term>sound source separation</term>
<term>α-stable</term>
</keywords>
</textClass>
</profileDesc>
</teiHeader>
<front><div type="abstract" xml:lang="en">We propose a projection-based method for the unmixing of multi-channel audio signals into their different constituent spatial objects. Here, spatial objects are modelled using a unified framework which handles both point sources and diffuse sources. We then propose a novel methodology to estimate and take advantage of the spatial dependencies of an object. Where previous research has processed the original multichannel mixtures directly and has been principally focused on the use of inter-channel covariance structures, here we instead process projections of the multichannel signal on many different spatial directions. These linear combinations consist of observations where some spatial objects are cancelled or enhanced. We then propose an algorithm which takes these projections as the observations, discarding dependencies between them. Since each one contains global information regarding all channels of the original multichannel mixture, this provides an effective means of learning the parameters of the original audio, while avoiding the need for joint-processing of all the channels. We further show how to recover the separated spatial objects and demonstrate the use of the technique on stereophonic music signals.</div>
</front>
</TEI>
<affiliations><list><country><li>France</li>
<li>Irlande (pays)</li>
</country>
<region><li>Grand Est</li>
<li>Lorraine (région)</li>
</region>
<settlement><li>Metz</li>
<li>Nancy</li>
</settlement>
<orgName><li>Université de Lorraine</li>
</orgName>
</list>
<tree><country name="Irlande (pays)"><noRegion><name sortKey="Fitzgerald, Derry" sort="Fitzgerald, Derry" uniqKey="Fitzgerald D" first="Derry" last="Fitzgerald">Derry Fitzgerald</name>
</noRegion>
</country>
<country name="France"><region name="Grand Est"><name sortKey="Liutkus, Antoine" sort="Liutkus, Antoine" uniqKey="Liutkus A" first="Antoine" last="Liutkus">Antoine Liutkus</name>
</region>
<name sortKey="Badeau, Roland" sort="Badeau, Roland" uniqKey="Badeau R" first="Roland" last="Badeau">Roland Badeau</name>
</country>
</tree>
</affiliations>
</record>
Pour manipuler ce document sous Unix (Dilib)
EXPLOR_STEP=$WICRI_ROOT/Wicri/Lorraine/explor/InforLorV4/Data/Main/Exploration
HfdSelect -h $EXPLOR_STEP/biblio.hfd -nk 000169 | SxmlIndent | more
Ou
HfdSelect -h $EXPLOR_AREA/Data/Main/Exploration/biblio.hfd -nk 000169 | SxmlIndent | more
Pour mettre un lien sur cette page dans le réseau Wicri
{{Explor lien |wiki= Wicri/Lorraine |area= InforLorV4 |flux= Main |étape= Exploration |type= RBID |clé= Hal:hal-01248014 |texte= PROJET - Spatial Audio Separation Using Projections }}
This area was generated with Dilib version V0.6.33. |